deepmind control suite experiment
A General Implementation Details For our Atari games and DeepMind Control Suite experiments, we largely follow DrQ [ 33
The replay buffer size is 100K. This procedure is repeated every time an image is sampled from the replay buffer. Batch size used in both RL and representation learning is 512. The corresponding hyperparameters used in Atari experiments are shown in Table 7 and Table 8. The action repeat hyperparameters are show in Table 6.